NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

DESCG: data encoding scheme classification with GNN in binary analysis

https://doi.org/10.1007/s10515-025-00538-0

Dai, Xushu; Luo, Nanqing; Wang, Haizhou; Wang, Zhilong; Cao, Chen; Liu, Peng (July 2025, Automated Software Engineering)

Abstract Binary analysis, the process of examining software without its source code, plays a crucial role in understanding program behavior, e.g., evaluating the security properties of commercial software, and analyzing malware. One challenging aspect of this process is to classify data encoding schemes, such as encryption and compression, due to the absence of high-level semantic information. Existing approaches either rely on code similarity, which only works for known schemes, or heuristic rules, which lack scalability. In this paper, we propose DESCG, a novel deep learning-based method for automatically classifying four widely employed kinds of data encoding schemes in binary programs: encryption, compression, decompression, and hashing. Our approach leverages dynamic analysis to extract execution traces from binary programs, builds data dependency graphs from these traces, and incorporates critical feature engineering. By combining the specialized graph representation with the Graph Neural Network (GNN), our approach enables accurate classification without requiring prior knowledge of specific encoding schemes. The Evaluation result shows that DESCG achieves 97.7% accuracy and an F1 score of 97.67%, outperforming baseline models. We also conducted an extensive evaluation of DESCG to explore which feature is more important for it and examine its performance and overhead.
more » « less
Free, publicly-accessible full text available July 18, 2026
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt

https://doi.org/10.1109/DSN-S65789.2025.00037

Wang, Zhilong; Nagaraja, Neha; Zhang, Lan; Bahsi, Hayretdin; Patil, Pawan; Liu, Peng (June 2025, IEEE)

Free, publicly-accessible full text available June 23, 2026
Position paper: GPT conjecture: understanding the trade-offs between granularity, performance and timeliness in control-flow integrity

https://doi.org/10.1186/s42400-021-00098-2

Wang, Zhilong; Liu, Peng (September 2021, Cybersecurity)

Abstract Performance/security trade-off is widely noticed in CFI research, however, we observe that not every CFI scheme is subject to the trade-off. Motivated by the key observation, we ask three questions: ➊ does trade-off really exist in different CFI schemes? ➋ if trade-off do exist, how do previous works comply with it? ➌ how can it inspire future research? Although the three questions probably cannot be directly answered, they are inspiring. We find that a deeper understanding of the nature of the trade-off will help answer the three questions. Accordingly, we proposed theGPTconjecture to pinpoint the trade-off in designing CFI schemes, which says that at most two out of three properties (fine granularity, acceptable performance, and preventive protection) could be achieved.
more » « less
Using deep learning to solve computer security challenges: a survey

https://doi.org/10.1186/s42400-020-00055-5

Choi, Yoon-Ho; Liu, Peng; Shang, Zitong; Wang, Haizhou; Wang, Zhilong; Zhang, Lan; Zhou, Junwei; Zou, Qingtian (December 2020, Cybersecurity)
null (Ed.)
Abstract Although using machine learning techniques to solve computer security challenges is not a new idea, the rapidly emerging Deep Learning technology has recently triggered a substantial amount of interests in the computer security community. This paper seeks to provide a dedicated review of the very recent research works on using Deep Learning techniques to solve computer security challenges. In particular, the review covers eight computer security problems being solved by applications of Deep Learning: security-oriented program analysis, defending return-oriented programming (ROP) attacks, achieving control-flow integrity (CFI), defending network attacks, malware classification, system-event-based anomaly detection, memory forensics, and fuzzing for software security.
more » « less
Full Text Available

Search for: All records